Papers with domain-specific fine-tuning of visual encoders

1 papers
An Efficient Gloss-Free Sign Language Translation Using Spatial Configurations and Motion Dynamics with LLMs (2025.naacl-long)

Copied to clipboard

Challenge: Existing methods for sign language translation rely on glosses, which are written representations of signs.
Approach: They propose a new LLM-based SLT framework that uses off-the-shelf visual encoders to extract spatial and motion features from sign videos.
Outcome: The proposed framework captures spatial configurations and motion dynamics in sign language without domain-specific tuning.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations